Papers by Camillo Jose Taylor
Towards Rationality in Language and Multimodal Agents: A Survey (2025.naacl-long)
Copied to clipboard
Bowen Jiang, Yangxinyu Xie, Xiaomeng Wang, Yuan Yuan, Zhuoqun Hao, Xinyi Bai, Weijie J Su, Camillo Jose Taylor, Tanwi Mallick
| Challenge: | despite advances in language and multimodal agents, large language models lack rationality . despite their progress, large-scale models lack real-world grounding and feedback mechanisms . |
| Approach: | They propose to build more rational language and multimodal agents . they also examine what criteria define rationality in intelligent systems . |
| Outcome: | This paper assesses the state-of-the-art in language and multimodal agents . it also outlines open challenges and future research directions . |
ControlText: Unlocking Controllable Fonts in Multilingual Text Rendering without Font Annotations (2025.findings-emnlp)
Copied to clipboard
Bowen Jiang, Yuan Yuan, Xinyi Bai, Zhuoqun Hao, Alyson Yin, Yaojie Hu, Wenyu Liao, Lyle Ungar, Camillo Jose Taylor
| Challenge: | a new method for visual text rendering requires glyph annotations to be obtained . |
| Approach: | They propose a model that integrates diffusion with a text segmentation model to achieve multilingual text rendering using just raw images without font label annotations. |
| Outcome: | The proposed model can achieve font-controllable multilingual text rendering without label annotations. |